Search CORE

29 research outputs found

Kernel Belief Propagation

Author: Bickson Danny
Gretton Arthur
Guestrin Carlos
Low Yucheng
Song Le
Publication venue
Publication date: 01/01/2011
Field of study

We propose a nonparametric generalization of belief propagation, Kernel Belief Propagation (KBP), for pairwise Markov random fields. Messages are represented as functions in a reproducing kernel Hilbert space (RKHS), and message updates are simple linear operations in the RKHS. KBP makes none of the assumptions commonly required in classical BP algorithms: the variables need not arise from a finite domain or a Gaussian distribution, nor must their relations take any particular parametric form. Rather, the relations between variables are represented implicitly, and are learned nonparametrically from training data. KBP has the advantage that it may be used on any domain where kernels are defined (Rd, strings, groups), even where explicit parametric models are not known, or closed form expressions for the BP updates do not exist. The computational cost of message updates in KBP is polynomial in the training data size. We also propose a constant time approximate message update procedure by representing messages using a small number of basis functions. In experiments, we apply KBP to image denoising, depth prediction from still images, and protein configuration prediction: KBP is faster than competing classical and nonparametric approaches (by orders of magnitude, in some cases), while providing significantly more accurate results

arXiv.org e-Print Archive

CiteSeerX

GraphLab: A New Framework for Parallel Machine Learning

Author: Bickson Danny
Gonzalez Joseph
Guestrin Carlos
Hellerstein Joseph M.
Kyrola Aapo
Low Yucheng
Publication venue
Publication date: 01/01/2010
Field of study

Designing and implementing efficient, provably correct parallel machine learning (ML) algorithms is challenging. Existing high-level parallel abstractions like MapReduce are insufficiently expressive while low-level tools like MPI and Pthreads leave ML experts repeatedly solving the same design challenges. By targeting common patterns in ML, we developed GraphLab, which improves upon abstractions like MapReduce by compactly expressing asynchronous iterative algorithms with sparse computational dependencies while ensuring data consistency and achieving a high degree of parallel performance. We demonstrate the expressiveness of the GraphLab framework by designing and implementing parallel versions of belief propagation, Gibbs sampling, Co-EM, Lasso and Compressed Sensing. We show that using GraphLab we can achieve excellent parallel performance on large scale real-world problems

arXiv.org e-Print Archive

CiteSeerX

Efficient Multicore Collaborative Filtering

Author: Bickson Danny
Low Yucheng
Wu Yao
Yan Qiang
Yang Qing
Publication venue
Publication date: 01/01/2011
Field of study

This paper describes the solution method taken by LeBuSiShu team for track1 in ACM KDD CUP 2011 contest (resulting in the 5th place). We identified two main challenges: the unique item taxonomy characteristics as well as the large data set size.To handle the item taxonomy, we present a novel method called Matrix Factorization Item Taxonomy Regularization (MFITR). MFITR obtained the 2nd best prediction result out of more then ten implemented algorithms. For rapidly computing multiple solutions of various algorithms, we have implemented an open source parallel collaborative filtering library on top of the GraphLab machine learning framework. We report some preliminary performance results obtained using the BlackLight supercomputer.Comment: In ACM KDD CUP Workshop 201

arXiv.org e-Print Archive

CiteSeerX

Distributed GraphLab: A Framework for Machine Learning in the Cloud

Author: Bickson Danny
Gonzalez Joseph
Guestrin Carlos
Hellerstein Joseph M.
Kyrola Aapo
Low Yucheng
Publication venue
Publication date: 01/01/2012
Field of study

While high-level data parallel frameworks, like MapReduce, simplify the design and implementation of large-scale data processing systems, they do not naturally or efficiently support many important data mining and machine learning algorithms and can lead to inefficient learning systems. To help fill this critical void, we introduced the GraphLab abstraction which naturally expresses asynchronous, dynamic, graph-parallel computation while ensuring data consistency and achieving a high degree of parallel performance in the shared-memory setting. In this paper, we extend the GraphLab framework to the substantially more challenging distributed setting while preserving strong data consistency guarantees. We develop graph based extensions to pipelined locking and data versioning to reduce network congestion and mitigate the effect of network latency. We also introduce fault tolerance to the GraphLab abstraction using the classic Chandy-Lamport snapshot algorithm and demonstrate how it can be easily implemented by exploiting the GraphLab abstraction itself. Finally, we evaluate our distributed implementation of the GraphLab abstraction on a large Amazon EC2 deployment and show 1-2 orders of magnitude performance gains over Hadoop-based implementations.Comment: VLDB201

arXiv.org e-Print Archive

CiteSeerX

Strong simulation: Capturing topology in graph pattern matching

Author: Abiteboul Serge
Amer-Yahia Sihem
Charu
Chia-Lung Chen Alan
Cormen Thomas H.
Dean Jeffrey
Diestel R.
Fard Arash
Gallagher Brian
Garey Michael
Giatsoglou Maria
Goldman Roy
Henzinger M. R.
Kann Viggo
Low Yucheng
Milner Robin
Papadimitriou Christos H
Publication venue: 'Association for Computing Machinery (ACM)'
Publication date: 01/01/2014
Field of study

Crossref

Edinburgh Research Explorer

Think Sequential, Run Parallel

Author: D Chazan
D Yan
DA Bader
DP Bertsekas
EP Xing
G Ramalingam
G. Ramalingam
GM Baudet
J Nesetril
Jeffrey Dean
K Andreev
L.G. VALIANT
LG Valiant
M Han
M Kim
Minyang Han
ML Fredman
Neil D. Jones
R. C. Prim
RG Gallager
W Fan
W Fan
Wenfei Fan
Yuanyuan Tian
Yucheng Low
Publication venue: 'Springer Science and Business Media LLC'
Publication date: 29/09/2018
Field of study

Crossref

Edinburgh Research Explorer